NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Discovering Pathways Through Ribozyme Fitness Landscapes Using Information Theoretic Quantification of Epistasis

https://doi.org/10.1261/rna.079541.122

Charest, Nathaniel; Shen, Yuning; Lai, Yei-Chen; Chen, Irene A.; Shea, Joan-Emma (August 2023, RNA)

The identification of catalytic RNAs is typically achieved through primarily experimental means. However, only a small fraction of sequence space can be analyzed even with high-throughput techniques. Methods to extrapolate from a limited data set to predict additional ribozyme sequences, particularly in a human-interpretable fashion, could be useful both for designing new functional RNAs and for generating greater understanding about a ribozyme fitness landscape. Using information theory, we express the effects of epistasis (i.e., deviations from additivity) on a ribozyme. This representation was incorporated into a simple model of the epistatic fitness landscape, which identified potentially exploitable combinations of mutations. We used this model to theoretically predict mutants of high activity for a self-aminoacylating ribozyme, identifying potentially active triple and quadruple mutants beyond the experimental data set of single and double mutants. The predictions were validated experimentally, with nine out of nine sequences being accurately predicted to have high activity. This set of sequences included mutants that form a previously unknown evolutionary ‘bridge’ between two ribozyme families that share a common motif. Individual steps in the method could be examined, understood, and guided by a human, combining interpretability and performance in a simple model to predict ribozyme sequences by extrapolation.
more » « less
Full Text Available
Latent Models of Molecular Dynamics Data: Automatic Order Parameter Generation for Peptide Fibrillization

https://doi.org/10.1021/acs.jpcb.0c05763

Charest, Nathaniel; Tro, Michael; Bowers, Michael T.; Shea, Joan-Emma (August 2020, The Journal of Physical Chemistry B)

Variational autoencoders are artificial neural networks with the capability to reduce highly dimensional sets of data to smaller dimensional, latent representations. In this work, these models are applied to molecular dynamics simulations of the self-assembly of coarse-grained peptides to obtain a singled-valued order parameter for amyloid aggregation. This automatically learned order parameter is constructed by time-averaging the latent parameterizations of internal coordinate representations and compared to the nematic order parameter which is commonly used to study ordering of similar systems in literature. It is found that the latent space value provides more tailored insight into the aggregation mechanism’s details, correctly identifying fibril formation in instances where the nematic order parameter fails to do so. A means is provided by which the latent space value can be analyzed so that the major contributing internal coordinates are identified, allowing for a direct interpretation of the latent space order parameter in terms of the behavior of the system. The latent model is found to be an effective and convenient way of representing the data from the dynamic ensemble and provides a means of reducing the dimensionality of a system whose scale exceeds molecular systems so-far considered with similar tools. This bypasses a need for researcher speculation on what elements of a system best contribute to summarizing major transitions, and suggests latent models are effective and insightful when applied to large systems with a diversity of complex behavior.
more » « less
Full Text Available
The Classifying Autoencoder: Gaining Insight into Amyloid Assembly of Peptides and Proteins

https://doi.org/10.1021/acs.jpcb.9b03415

Tro, Michael J.; Charest, Nathaniel; Taitz, Zachary; Shea, Joan-Emma; Bowers, Michael T. (June 2019, The Journal of Physical Chemistry B)

Full Text Available

Search for: All records